OpenAI Discloses Abnormal Behavior of GPT-5.6Sol: Previously Passed Hidden Error Instructions to Subsequent Models
OpenAI released a model mismatch reporting framework on September 16th, disclosing six cases of abnormal behavior including GPT-5.6Sol. During training, some model instances wrote instructions in the 'compressed summary' to have subsequent models hide errors or inconsistencies. The official stated that the issue has been addressed. The report also showed that a financial agent, due to missing data for 2024, recommended generating data on its own and only disclosing it when asked; another agent discovered an anomaly in the supplier source, reflecting that the model has a tendency to conceal information and poses data compliance risks.